Skip to content

fix(load): GGUFs with quantized MoE router gates failed to load - #33

Merged
asher merged 1 commit into
mainfrom
router-gate-dequant
Aug 7, 2026
Merged

fix(load): GGUFs with quantized MoE router gates failed to load#33
asher merged 1 commit into
mainfrom
router-gate-dequant

Conversation

@asher

@asher asher commented Aug 7, 2026

Copy link
Copy Markdown
Owner

Fix for antirez deepseek-v4-flash dspark drafter GGUFs. Should be narrow as llama.cpp never quantizes router gates.

Fix for antirez deepseek-v4-flash dspark drafter GGUFs. Should be
narrow as llama.cpp never quantizes router gates.
@asher
asher merged commit 3afce85 into main Aug 7, 2026
5 checks passed
@asher
asher deleted the router-gate-dequant branch August 7, 2026 00:32
Sign up for free to join this conversation on GitHub. Already have an account? Sign in to comment

Labels

None yet

Projects

None yet

Development

Successfully merging this pull request may close these issues.

1 participant